‹ BackNewsOpus 5

Opus 5

Google launches Gemini 3.8 Flash at unchanged pricing, with benchmark wins over Opus 5
Anthropic’s Fable 5.1 reportedly enters limited rollout as launch speculation builds
Claude test leaks point to strong 3D reasoning in two unreleased models
FT: Anthropic’s Fable 5 still holds only about 11% of enterprise spend more than two months after launch
Anthropic
2026-08-24 01:14:10

Anthropic faces Claude Code backlash after hidden effort-mapping test sparks downgrade claims

Anthropic has come under fire after developers said Claude Code appeared to get worse without notice, only for one engineer to trace the issue to a hidden experiment in how the product mapped reasoning effort values. Developer argofowl said he spent an afternoon debugging what looked like broken behavior, eventually finding that requests marked as “high” in Claude Code were showing up as “10” in the API logs. He said the change was not mentioned in the product’s changelog. According to the discussion cited in the source material, the behavior appeared in Claude Code 2.1.237 and was tied to an experiment affecting Fable 5 sessions on version 2.1.236 and later. Older versions and Opus 5 were said to be unaffected by that specific test. Claude Code engineer Thariq Shihipar replied that Anthropic sometimes tests API service configurations inside Claude Code before deciding whether to roll them out broadly. He said the live experiment only changed the numeric mapping for effort and that the number itself should not be read on a 0-to-100 scale. In his words, users still received the effort level they selected, and internal evaluations found no performance impact. Even as that explanation addressed one part of the uproar, a separate wave of complaints hit Opus 5, which some users described as unstable, lazy, and error-prone. Shihipar later acknowledged publicly that Opus 5’s performance was inconsistent and said fixing it had become a top priority for the team.

1140
Anthropic faces Claude Code backlash after hidden effort-mapping test sparks downgrade claims
Leaked Claude model IDs put Anthropic’s Fable 5 adoption problem back in focus
Opus 5 jumps from 30.2% to 96.2% on ARC-AGI-3 in Jeremy Berman’s computer-based test